Papers with Online learning
SaFeRDialogues: Taking Feedback Gracefully after Conversational Safety Failures (2022.acl-long)
Copied to clipboard
| Challenge: | Existing open-domain conversational models can easily be made to talk in inadequate ways. |
| Approach: | They propose a task and dataset of graceful responses to safety feedback . they collect 8k dialogues demonstrating safety failures, feedback signaling them, and a response acknowledging feedback. |
| Outcome: | The proposed model improves on a dataset of 8k dialogues demonstrating safety failures, feedback signaling them, and a response acknowledging the feedback. |